Papers by Yong Keong Yap
Improving the Detection of Multilingual Online Attacks with Rich Social Media Data from Singapore (2023.acl-long)
Copied to clipboard
Janosch Haber, Bertie Vidgen, Matthew Chapman, Vibhor Agarwal, Roy Ka-Wei Lee, Yong Keong Yap, Paul Röttger
| Challenge: | Toxic content is a global problem, but most resources for detecting toxic content are in English . new datasets and models for non-English languages focus exclusively on one language or dialect . |
| Approach: | They propose to use a multilingual dataset of online attacks to identify code-mixed toxic content in Singapore . they collect reddit comments in Indonesian, Malay, Singlish, and other languages and provide fine-grained hierarchical labels for attacks . |
| Outcome: | The proposed dataset provides fine-grained hierarchical labels for online attacks in Singapore . it shows that the metadata can be used for granular error analysis . |
Guiding Computational Stance Detection with Expanded Stance Triangle Framework (2023.acl-long)
Copied to clipboard
| Challenge: | Experimental results show that strategically-enriched data can significantly improve the performance on out-of-domain and cross-target evaluation. |
| Approach: | They propose to decompose a stance detection task from a theoretical perspective and extend it with additional annotations. |
| Outcome: | The proposed task improves performance on out-of-domain and cross-target evaluations using a linguistic framework. |